AITopics | residual network behave

Collaborating Authors

residual network behave

Information about AI from the News, Publications, and Conferences

Automatic Classification – Tagging and Summarization – Customizable Filtering and Analysis

If you are looking for an answer to the question What is Artificial Intelligence? and you only have a minute, then here's the definition the Association for the Advancement of Artificial Intelligence offers on its home page: "the scientific understanding of the mechanisms underlying thought and intelligent behavior and their embodiment in machines."

However, if you are fortunate enough to have more than a minute, then please get ready to embark upon an exciting journey exploring AI (but beware, it could last a lifetime) …

Residual Networks Behave Like Ensembles of Relatively Shallow Networks

Neural Information Processing SystemsNov-21-2025, 14:32:34 GMT

In this work we propose a novel interpretation of residual networks showing that they can be seen as a collection of many paths of differing length. Moreover, residual networks seem to enable very deep networks by leveraging only the short paths during training. To support this observation, we rewrite residual networks as an explicit collection of paths. Unlike traditional models, paths through residual networks vary in length. Further, a lesion study reveals that these paths show ensemble-like behavior in the sense that they do not strongly depend on each other. Finally, and most surprising, most paths are shorter than one might expect, and only the short paths are needed during training, as longer paths do not contribute any gradient. For example, most of the gradient in a residual network with 110 layers comes from paths that are only 10-34 layers deep. Our results reveal one of the key characteristics that seem to enable the training of very deep networks: Residual networks avoid the vanishing gradient problem by introducing short paths which can carry gradient throughout the extent of very deep networks.

ensemble, name change, residual network behave, (5 more...)

Neural Information Processing Systems

Technology: Information Technology > Artificial Intelligence > Machine Learning (0.40)

Add feedback

Reviews: Residual Networks Behave Like Ensembles of Relatively Shallow Networks

Neural Information Processing SystemsJan-20-2025, 09:41:50 GMT

I think the experiments in this paper are very interesting and reveal interesting properties of Resnet. However, I don't think the claim that ResNets are "exponential ensembles" has been made adequetely. If the claim were weakened to "behave like exponential ensembles", this would seem to be reasonably well supported by the experiments, and it's also an interesting result. But I think the question of why this happens is still open. One thing that stands out is that no definition of "ensemble" is ever given, not even a rough definition.

ensemble, exponential ensemble, residual network behave, (6 more...)

Neural Information Processing Systems

Technology: Information Technology > Artificial Intelligence > Machine Learning (0.35)

Add feedback